Papers with multi-faceted analysis
Cascade versus Direct Speech Translation: Do the Differences Still Make a Difference? (2021.acl-long)
Copied to clipboard
Luisa Bentivogli, Mauro Cettolo, Marco Gaido, Alina Karakanta, Alberto Martinelli, Matteo Negri, Marco Turchi
| Challenge: | a gap between direct approaches to speech translation (ST) and traditional cascade solutions has gradually decreased . a recent study found that the subtle differences observed in their behavior are not sufficient for humans neither to distinguish them nor to prefer one over the other. |
| Approach: | They compare state-of-the-art systems representative of the two paradigms . they find subtle differences observed in their behavior are not sufficient . |
| Outcome: | The proposed system is compared with state-of-the-art systems representative of the two paradigms. |
Unveiling Intrinsic Dimension of Texts: from Academic Abstract to Creative Story (2026.eacl-long)
Copied to clipboard
Pedashenko Vladislav, Laida Kushnareva, Yana Khassan Nibal, Eduard Tulchinskii, Kristian Kuznetsov, Vladislav Zharchinskii, Yury Maximov, Irina Piontkovskaya
| Challenge: | a new study grounding intrinsic dimension in interpretable text properties is published . entropy-like measures are ubiquitous in training and evaluation, but geometric complexity remains underexplored. |
| Approach: | They propose to ground intrinsic dimension (ID) in interpretable text properties through cross-encoder analysis, linguistic features, and sparse autoencodes. |
| Outcome: | The proposed method shows that scientific prose shows low ID ( 8), encyclopedic content medium ID ( > 9), creative/opinion writing high ID (> 10.5) |
Beyond Static Benchmarks: Synthesizing Harmful Content via Persona-based Simulation for Robust Evaluation (2026.acl-long)
Copied to clipboard
| Challenge: | Existing static benchmarks for harmful content detection face limitations in scalability and diversity. |
| Approach: | They propose a framework for synthesizing harmful content using persona-guided large language model agents. |
| Outcome: | The proposed framework achieves a high success rate in harmful generation tests across multiple detection systems. |